AntGroup Open Sources SingProbe: Real-Time Risk Blocking for Large Model Generation with Lightweight Probe, Computing Overhead Below 0.5%
Ant Group's AI Security Lab opens the internal security guardrail technology called SingProbe. Unlike external review, it reuses the hidden states of the base model for inference, achieving real-time token-level risk assessment with less than 0.5% additional computing power. It can simultaneously perform intent classification, safety detection, and hallucination recognition, and can block risks at the millisecond level before output in high-risk scenarios such as healthcare. It has been adapted to 29 mainstream large models.